High dimensional data analysis using multivariate generalized spatial quantiles

نویسندگان

  • Nitai D. Mukhopadhyay
  • Snigdhansu Chatterjee
چکیده

High dimensional data routinely arises in image analysis, genetic experiments, network analysis, and various other research areas. Many such datasets do not correspond to well-studied probability distributions, and in several applications the data-cloud prominently displays non-symmetric and non-convex shape features. We propose using spatial quantiles and their generalizations, in particular, the projection quantile, for describing, analyzing and conducting inference with multivariate data. Minimal assumptions are made about the nature and shape characteristics of the underlying probability distribution, and we do not require the sample size to be as high as the data-dimension. We present theoretical properties of the generalized spatial quantiles, and an algorithm to compute them quickly. Our quantiles may be used to obtain multidimensional confidence or credible regions that are not required to conform to a pre-determined shape. We also propose a new notion of multidimensional order statistics, which may be used to obtain multidimensional outliers. Many of the features revealed using a generalized spatial quantile-based analysis would be missed if the data was shoehorned into a well-known probabilistic configuration.

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Topological and canonical kriging for design flood prediction in ungauged catchments: an improvement over a traditional regional regression approach?

In the United States, estimation of flood frequency quantiles at ungauged locations has been largely based on regional regression techniques that relate measurable catchment descriptors to flood quantiles. More recently, spatial interpolation techniques of point data have been shown to be effective for predicting streamflow statistics (i.e., flood flows and low-flow indices) in ungauged catchme...

متن کامل

Nonparametric Multivariate Descriptive Measures Based on Spatial Quantiles

An appealing way of working with probability distributions, especially in nonparametric inference, is through “descriptive measures” that characterize features of particular interest. One attractive approach is to base the measures on quantiles. Here we consider the multivariate context and utilize the “spatial quantiles”, a recent vector extension of univariate quantiles that is becoming incre...

متن کامل

Factorisable Sparse Tail Event Curves

In this paper, we propose a multivariate quantile regression method which enables localized analysis on conditional quantiles and global comovement analysis on conditional ranges for high-dimensional data. The proposed method, hereafter referred to as FActorisable Sparse Tail Event Curves, or FASTEC for short, exploits the potential factor structure of multivariate conditional quantiles through...

متن کامل

Generalized Multivariate Rank Type Test Statistics via Spatial U-Quantiles

The classical univariate sign and signed rank tests for location have been extended over the years to the multivariate setting, including recent robust rotation invariant “spatial” versions. Here we introduce a broad class of rotation invariant multivariate spatial generalized rank type test statistics. For a given inference problem not restricted to location, the test statistics are linked thr...

متن کامل

Bahadur representation for U-quantiles of dependent data

Abstract. U -quantiles are applied in robust statistics, like the Hodges-Lehmann estimator of location for example. They have been analyzed in the case of independent random variables with the help of a generalized Bahadur representation. Our main aim is to extend these results to U -quantiles of strongly mixing random variables and functionals of absolutely regular sequences. We obtain the cen...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

عنوان ژورنال:
  • Journal of multivariate analysis

دوره 102 4  شماره 

صفحات  -

تاریخ انتشار 2011